The AI arena is free today

Open Superagent

Graphwalks parents >128k

Progress Over Time

Interactive timeline showing model performance evolution on Graphwalks parents >128k

State-of-the-art frontier
Open
Proprietary

Graphwalks parents >128k Leaderboard

7 models
ContextCostLicense
1—1.0M$5.00 / $25.00
2—1.0M$5.00 / $25.00
3
OpenAI
OpenAI
—1.1M$5.00 / $30.00
4
OpenAI
OpenAI
—1.0M$2.50 / $15.00
5
OpenAI
OpenAI
—1.0M$2.00 / $8.00
6—1.0M$0.40 / $1.60
7—1.0M$0.10 / $0.40
Notice missing or incorrect data?
About this benchmark

What is Graphwalks parents >128k?

A graph reasoning benchmark that evaluates language models' ability to find parent nodes in graphs with context length over 128k tokens, testing long-context reasoning and graph structure understanding.

Graphwalks parents >128k is a text benchmark evaluating models on long context, reasoning, and spatial reasoning tasks. LLM Stats tracks 7 models on this benchmark, scored on a 0–1 scale. The current average is 0.4, with the leader at 1.0.

Compare leaders on the best AI for long context, best AI for reasoning and best AI for spatial reasoning leaderboards.

Current leaders

Claude Opus 4.6 from Anthropic currently leads the Graphwalks parents >128k leaderboard with a score of 0.954 across 7 evaluated AI models.

1Claude Opus 4.6Anthropic95.4%
2Claude Opus 4.8Anthropic83.3%
3GPT-5.5OpenAI58.5%

FAQ

Common questions about the Graphwalks parents >128k benchmark and leaderboard.

What is the Graphwalks parents >128k benchmark?

A graph reasoning benchmark that evaluates language models' ability to find parent nodes in graphs with context length over 128k tokens, testing long-context reasoning and graph structure understanding.

What is the Graphwalks parents >128k leaderboard?

The Graphwalks parents >128k leaderboard ranks 7 AI models based on their performance on this benchmark. Currently, Claude Opus 4.6 by Anthropic leads with a score of 0.954. The average score across all models is 0.445.

What is the highest Graphwalks parents >128k score?

The highest Graphwalks parents >128k score is 0.954, achieved by Claude Opus 4.6 from Anthropic.

How many models are evaluated on Graphwalks parents >128k?

7 models have been evaluated on the Graphwalks parents >128k benchmark, with 0 verified results and 7 self-reported results.

What categories does Graphwalks parents >128k cover?

Graphwalks parents >128k is categorized under long context, reasoning, and spatial reasoning. The benchmark evaluates text models.

Which model offers the best value on Graphwalks parents >128k?

Among models scoring within 10% of the leader, Claude Opus 4.6 from Anthropic is the cheapest, at $5.00 per million input tokens with a score of 0.954.

How recent are the Graphwalks parents >128k leaderboard results?

The Graphwalks parents >128k leaderboard was last updated in October 2026 and currently includes 7 evaluated models.
Graphwalks parents >128k Leaderboard